Serveur d'exploration sur la recherche en informatique en Lorraine

Attention, ce site est en cours de développement !
Attention, site généré par des moyens informatiques à partir de corpus bruts.
Les informations ne sont donc pas validées.

Real Time Context-Independent Phone Recognition Using a Simplified Statistical Training Algorithm

Identifieur interne : 001B49 ( Main/Exploration ); précédent : 001B48; suivant : 001B50

Real Time Context-Independent Phone Recognition Using a Simplified Statistical Training Algorithm

Auteurs : Othman Lachhab [Maroc] ; Joseph Di Martino [France] ; El Hassan Ibn Elhaj [Maroc] ; Ahmed Hammouch [France]

Source :

RBID : Hal:hal-00761816

English descriptors

Abstract

In this paper we present our own real time speaker-independent continuous phone recognition (Spirit) using Context-Independent Continuous Density HMMs (CI-CDHMMs) modeled by Gaussian Mixtures Models (GMMs). All the parameters of our system are estimated directly from data by using an improved Viterbi alignment process instead of the classical Baum-Welch estimation procedure. Generally, in the literature the Viterbi training algorithm is used as a pretreatment to initialize HMMs models that will be most often re-estimated by using complex re-estimation formula. In order to evaluate and compare the performance of our system with other previous works, we use the TIMIT database. The duration test of our recognition system for each sentence is between 2 seconds (for short sentences) to 12 seconds (for long sentences). We get, by combining the 64 possible phones into 39 phonetic classes, a phone recognition correct rate of 71.06% and an accuracy rate of 65.25%. These results compare favorably with previously published works.

Url:


Affiliations:


Links toward previous steps (curation, corpus...)


Le document en format XML

<record>
<TEI>
<teiHeader>
<fileDesc>
<titleStmt>
<title xml:lang="en">Real Time Context-Independent Phone Recognition Using a Simplified Statistical Training Algorithm</title>
<author>
<name sortKey="Lachhab, Othman" sort="Lachhab, Othman" uniqKey="Lachhab O" first="Othman" last="Lachhab">Othman Lachhab</name>
<affiliation wicri:level="1">
<hal:affiliation type="institution" xml:id="struct-332892" status="VALID">
<orgName>Institut National de Postes et Télécommunications [Rabat]</orgName>
<orgName type="acronym">INPT</orgName>
<desc>
<address>
<addrLine>2, av ALLal EL Fassi - Madinat AL Irfane - Rabat</addrLine>
<country key="MA"></country>
</address>
<ref type="url">http://www.inpt.ac.ma</ref>
</desc>
</hal:affiliation>
<country>Maroc</country>
</affiliation>
</author>
<author>
<name sortKey="Di Martino, Joseph" sort="Di Martino, Joseph" uniqKey="Di Martino J" first="Joseph" last="Di Martino">Joseph Di Martino</name>
<affiliation wicri:level="1">
<hal:affiliation type="researchteam" xml:id="struct-2359" status="OLD">
<idno type="RNSR">200118295L</idno>
<orgName>Analysis, perception and recognition of speech</orgName>
<orgName type="acronym">PAROLE</orgName>
<desc>
<address>
<country key="FR"></country>
</address>
<ref type="url">http://www.inria.fr/equipes/parole</ref>
</desc>
<listRelation>
<relation active="#struct-160" type="direct"></relation>
<relation name="UMR7503" active="#struct-441569" type="indirect"></relation>
<relation active="#struct-300009" type="indirect"></relation>
<relation active="#struct-300291" type="indirect"></relation>
<relation active="#struct-300292" type="indirect"></relation>
<relation active="#struct-300293" type="indirect"></relation>
<relation active="#struct-2496" type="direct"></relation>
</listRelation>
<tutelles>
<tutelle active="#struct-160" type="direct">
<org type="laboratory" xml:id="struct-160" status="OLD">
<orgName>Laboratoire Lorrain de Recherche en Informatique et ses Applications</orgName>
<orgName type="acronym">LORIA</orgName>
<desc>
<address>
<addrLine>Campus Scientifique BP 239 54506 Vandoeuvre-lès-Nancy Cedex</addrLine>
<country key="FR"></country>
</address>
<ref type="url">http://www.loria.fr</ref>
</desc>
<listRelation>
<relation name="UMR7503" active="#struct-441569" type="direct"></relation>
<relation active="#struct-300009" type="direct"></relation>
<relation active="#struct-300291" type="direct"></relation>
<relation active="#struct-300292" type="direct"></relation>
<relation active="#struct-300293" type="direct"></relation>
</listRelation>
</org>
</tutelle>
<tutelle name="UMR7503" active="#struct-441569" type="indirect">
<org type="institution" xml:id="struct-441569" status="VALID">
<idno type="IdRef">02636817X</idno>
<idno type="ISNI">0000000122597504</idno>
<orgName>Centre National de la Recherche Scientifique</orgName>
<orgName type="acronym">CNRS</orgName>
<date type="start">1939-10-19</date>
<desc>
<address>
<country key="FR"></country>
</address>
<ref type="url">http://www.cnrs.fr/</ref>
</desc>
</org>
</tutelle>
<tutelle active="#struct-300009" type="indirect">
<org type="institution" xml:id="struct-300009" status="VALID">
<orgName>Institut National de Recherche en Informatique et en Automatique</orgName>
<orgName type="acronym">Inria</orgName>
<desc>
<address>
<addrLine>Domaine de VoluceauRocquencourt - BP 10578153 Le Chesnay Cedex</addrLine>
<country key="FR"></country>
</address>
<ref type="url">http://www.inria.fr/en/</ref>
</desc>
</org>
</tutelle>
<tutelle active="#struct-300291" type="indirect">
<org type="institution" xml:id="struct-300291" status="OLD">
<orgName>Université Henri Poincaré - Nancy 1</orgName>
<orgName type="acronym">UHP</orgName>
<date type="end">2011-12-31</date>
<desc>
<address>
<addrLine>24-30 rue Lionnois, BP 60120, 54 003 NANCY cedex, France</addrLine>
<country key="FR"></country>
</address>
</desc>
</org>
</tutelle>
<tutelle active="#struct-300292" type="indirect">
<org type="institution" xml:id="struct-300292" status="OLD">
<orgName>Université Nancy 2</orgName>
<date type="end">2011-12-31</date>
<desc>
<address>
<addrLine>91 avenue de la Libération, BP 454, 54001 Nancy cedex</addrLine>
<country key="FR"></country>
</address>
</desc>
</org>
</tutelle>
<tutelle active="#struct-300293" type="indirect">
<org type="institution" xml:id="struct-300293" status="OLD">
<orgName>Institut National Polytechnique de Lorraine</orgName>
<orgName type="acronym">INPL</orgName>
<date type="end">2011-12-31</date>
<desc>
<address>
<country key="FR"></country>
</address>
</desc>
</org>
</tutelle>
<tutelle active="#struct-2496" type="direct">
<org type="laboratory" xml:id="struct-2496" status="OLD">
<orgName>INRIA Lorraine</orgName>
<desc>
<address>
<addrLine>615 rue du Jardin Botanique 54600 Villers-lès-Nancy</addrLine>
<country key="FR"></country>
</address>
<ref type="url">http://www.inria.fr/centre-de-recherche-inria/nancy-grand-est</ref>
</desc>
<listRelation>
<relation active="#struct-300009" type="direct"></relation>
</listRelation>
</org>
</tutelle>
</tutelles>
</hal:affiliation>
<country>France</country>
<placeName>
<settlement type="city">Nancy</settlement>
<region type="region" nuts="2">Grand Est</region>
<region type="old region" nuts="2">Lorraine (région)</region>
</placeName>
<orgName type="university">Université Nancy 2</orgName>
<orgName type="institution" wicri:auto="newGroup">Université de Lorraine</orgName>
<placeName>
<settlement type="city">Nancy</settlement>
<region type="region" nuts="2">Grand Est</region>
<region type="old region" nuts="2">Lorraine (région)</region>
</placeName>
<orgName type="university">Institut national polytechnique de Lorraine</orgName>
<orgName type="institution" wicri:auto="newGroup">Université de Lorraine</orgName>
</affiliation>
</author>
<author>
<name sortKey="Ibn Elhaj, El Hassan" sort="Ibn Elhaj, El Hassan" uniqKey="Ibn Elhaj E" first="El Hassan" last="Ibn Elhaj">El Hassan Ibn Elhaj</name>
<affiliation wicri:level="1">
<hal:affiliation type="institution" xml:id="struct-332892" status="VALID">
<orgName>Institut National de Postes et Télécommunications [Rabat]</orgName>
<orgName type="acronym">INPT</orgName>
<desc>
<address>
<addrLine>2, av ALLal EL Fassi - Madinat AL Irfane - Rabat</addrLine>
<country key="MA"></country>
</address>
<ref type="url">http://www.inpt.ac.ma</ref>
</desc>
</hal:affiliation>
<country>Maroc</country>
</affiliation>
</author>
<author>
<name sortKey="Hammouch, Ahmed" sort="Hammouch, Ahmed" uniqKey="Hammouch A" first="Ahmed" last="Hammouch">Ahmed Hammouch</name>
<affiliation wicri:level="1">
<hal:affiliation type="laboratory" xml:id="struct-208694" status="INCOMING">
<orgName>Ecole Normale Supérieure de l'Enseignement Technique</orgName>
<orgName type="acronym">ENSET</orgName>
<desc>
<address>
<country key="FR"></country>
</address>
</desc>
<listRelation>
<relation active="#struct-351043" type="direct"></relation>
</listRelation>
<tutelles>
<tutelle active="#struct-351043" type="direct">
<org type="institution" xml:id="struct-351043" status="INCOMING">
<orgName>Ecole normale supérieure de l'Enseignement Technique</orgName>
<desc>
<address>
<country key="FR"></country>
</address>
</desc>
</org>
</tutelle>
</tutelles>
</hal:affiliation>
<country>France</country>
</affiliation>
</author>
</titleStmt>
<publicationStmt>
<idno type="wicri:source">HAL</idno>
<idno type="RBID">Hal:hal-00761816</idno>
<idno type="halId">hal-00761816</idno>
<idno type="halUri">https://hal.inria.fr/hal-00761816</idno>
<idno type="url">https://hal.inria.fr/hal-00761816</idno>
<date when="2012-05-10">2012-05-10</date>
<idno type="wicri:Area/Hal/Corpus">003F64</idno>
<idno type="wicri:Area/Hal/Curation">003F64</idno>
<idno type="wicri:Area/Hal/Checkpoint">001769</idno>
<idno type="wicri:explorRef" wicri:stream="Hal" wicri:step="Checkpoint">001769</idno>
<idno type="wicri:Area/Main/Merge">001B79</idno>
<idno type="wicri:Area/Main/Curation">001B49</idno>
<idno type="wicri:Area/Main/Exploration">001B49</idno>
</publicationStmt>
<sourceDesc>
<biblStruct>
<analytic>
<title xml:lang="en">Real Time Context-Independent Phone Recognition Using a Simplified Statistical Training Algorithm</title>
<author>
<name sortKey="Lachhab, Othman" sort="Lachhab, Othman" uniqKey="Lachhab O" first="Othman" last="Lachhab">Othman Lachhab</name>
<affiliation wicri:level="1">
<hal:affiliation type="institution" xml:id="struct-332892" status="VALID">
<orgName>Institut National de Postes et Télécommunications [Rabat]</orgName>
<orgName type="acronym">INPT</orgName>
<desc>
<address>
<addrLine>2, av ALLal EL Fassi - Madinat AL Irfane - Rabat</addrLine>
<country key="MA"></country>
</address>
<ref type="url">http://www.inpt.ac.ma</ref>
</desc>
</hal:affiliation>
<country>Maroc</country>
</affiliation>
</author>
<author>
<name sortKey="Di Martino, Joseph" sort="Di Martino, Joseph" uniqKey="Di Martino J" first="Joseph" last="Di Martino">Joseph Di Martino</name>
<affiliation wicri:level="1">
<hal:affiliation type="researchteam" xml:id="struct-2359" status="OLD">
<idno type="RNSR">200118295L</idno>
<orgName>Analysis, perception and recognition of speech</orgName>
<orgName type="acronym">PAROLE</orgName>
<desc>
<address>
<country key="FR"></country>
</address>
<ref type="url">http://www.inria.fr/equipes/parole</ref>
</desc>
<listRelation>
<relation active="#struct-160" type="direct"></relation>
<relation name="UMR7503" active="#struct-441569" type="indirect"></relation>
<relation active="#struct-300009" type="indirect"></relation>
<relation active="#struct-300291" type="indirect"></relation>
<relation active="#struct-300292" type="indirect"></relation>
<relation active="#struct-300293" type="indirect"></relation>
<relation active="#struct-2496" type="direct"></relation>
</listRelation>
<tutelles>
<tutelle active="#struct-160" type="direct">
<org type="laboratory" xml:id="struct-160" status="OLD">
<orgName>Laboratoire Lorrain de Recherche en Informatique et ses Applications</orgName>
<orgName type="acronym">LORIA</orgName>
<desc>
<address>
<addrLine>Campus Scientifique BP 239 54506 Vandoeuvre-lès-Nancy Cedex</addrLine>
<country key="FR"></country>
</address>
<ref type="url">http://www.loria.fr</ref>
</desc>
<listRelation>
<relation name="UMR7503" active="#struct-441569" type="direct"></relation>
<relation active="#struct-300009" type="direct"></relation>
<relation active="#struct-300291" type="direct"></relation>
<relation active="#struct-300292" type="direct"></relation>
<relation active="#struct-300293" type="direct"></relation>
</listRelation>
</org>
</tutelle>
<tutelle name="UMR7503" active="#struct-441569" type="indirect">
<org type="institution" xml:id="struct-441569" status="VALID">
<idno type="IdRef">02636817X</idno>
<idno type="ISNI">0000000122597504</idno>
<orgName>Centre National de la Recherche Scientifique</orgName>
<orgName type="acronym">CNRS</orgName>
<date type="start">1939-10-19</date>
<desc>
<address>
<country key="FR"></country>
</address>
<ref type="url">http://www.cnrs.fr/</ref>
</desc>
</org>
</tutelle>
<tutelle active="#struct-300009" type="indirect">
<org type="institution" xml:id="struct-300009" status="VALID">
<orgName>Institut National de Recherche en Informatique et en Automatique</orgName>
<orgName type="acronym">Inria</orgName>
<desc>
<address>
<addrLine>Domaine de VoluceauRocquencourt - BP 10578153 Le Chesnay Cedex</addrLine>
<country key="FR"></country>
</address>
<ref type="url">http://www.inria.fr/en/</ref>
</desc>
</org>
</tutelle>
<tutelle active="#struct-300291" type="indirect">
<org type="institution" xml:id="struct-300291" status="OLD">
<orgName>Université Henri Poincaré - Nancy 1</orgName>
<orgName type="acronym">UHP</orgName>
<date type="end">2011-12-31</date>
<desc>
<address>
<addrLine>24-30 rue Lionnois, BP 60120, 54 003 NANCY cedex, France</addrLine>
<country key="FR"></country>
</address>
</desc>
</org>
</tutelle>
<tutelle active="#struct-300292" type="indirect">
<org type="institution" xml:id="struct-300292" status="OLD">
<orgName>Université Nancy 2</orgName>
<date type="end">2011-12-31</date>
<desc>
<address>
<addrLine>91 avenue de la Libération, BP 454, 54001 Nancy cedex</addrLine>
<country key="FR"></country>
</address>
</desc>
</org>
</tutelle>
<tutelle active="#struct-300293" type="indirect">
<org type="institution" xml:id="struct-300293" status="OLD">
<orgName>Institut National Polytechnique de Lorraine</orgName>
<orgName type="acronym">INPL</orgName>
<date type="end">2011-12-31</date>
<desc>
<address>
<country key="FR"></country>
</address>
</desc>
</org>
</tutelle>
<tutelle active="#struct-2496" type="direct">
<org type="laboratory" xml:id="struct-2496" status="OLD">
<orgName>INRIA Lorraine</orgName>
<desc>
<address>
<addrLine>615 rue du Jardin Botanique 54600 Villers-lès-Nancy</addrLine>
<country key="FR"></country>
</address>
<ref type="url">http://www.inria.fr/centre-de-recherche-inria/nancy-grand-est</ref>
</desc>
<listRelation>
<relation active="#struct-300009" type="direct"></relation>
</listRelation>
</org>
</tutelle>
</tutelles>
</hal:affiliation>
<country>France</country>
<placeName>
<settlement type="city">Nancy</settlement>
<region type="region" nuts="2">Grand Est</region>
<region type="old region" nuts="2">Lorraine (région)</region>
</placeName>
<orgName type="university">Université Nancy 2</orgName>
<orgName type="institution" wicri:auto="newGroup">Université de Lorraine</orgName>
<placeName>
<settlement type="city">Nancy</settlement>
<region type="region" nuts="2">Grand Est</region>
<region type="old region" nuts="2">Lorraine (région)</region>
</placeName>
<orgName type="university">Institut national polytechnique de Lorraine</orgName>
<orgName type="institution" wicri:auto="newGroup">Université de Lorraine</orgName>
</affiliation>
</author>
<author>
<name sortKey="Ibn Elhaj, El Hassan" sort="Ibn Elhaj, El Hassan" uniqKey="Ibn Elhaj E" first="El Hassan" last="Ibn Elhaj">El Hassan Ibn Elhaj</name>
<affiliation wicri:level="1">
<hal:affiliation type="institution" xml:id="struct-332892" status="VALID">
<orgName>Institut National de Postes et Télécommunications [Rabat]</orgName>
<orgName type="acronym">INPT</orgName>
<desc>
<address>
<addrLine>2, av ALLal EL Fassi - Madinat AL Irfane - Rabat</addrLine>
<country key="MA"></country>
</address>
<ref type="url">http://www.inpt.ac.ma</ref>
</desc>
</hal:affiliation>
<country>Maroc</country>
</affiliation>
</author>
<author>
<name sortKey="Hammouch, Ahmed" sort="Hammouch, Ahmed" uniqKey="Hammouch A" first="Ahmed" last="Hammouch">Ahmed Hammouch</name>
<affiliation wicri:level="1">
<hal:affiliation type="laboratory" xml:id="struct-208694" status="INCOMING">
<orgName>Ecole Normale Supérieure de l'Enseignement Technique</orgName>
<orgName type="acronym">ENSET</orgName>
<desc>
<address>
<country key="FR"></country>
</address>
</desc>
<listRelation>
<relation active="#struct-351043" type="direct"></relation>
</listRelation>
<tutelles>
<tutelle active="#struct-351043" type="direct">
<org type="institution" xml:id="struct-351043" status="INCOMING">
<orgName>Ecole normale supérieure de l'Enseignement Technique</orgName>
<desc>
<address>
<country key="FR"></country>
</address>
</desc>
</org>
</tutelle>
</tutelles>
</hal:affiliation>
<country>France</country>
</affiliation>
</author>
</analytic>
</biblStruct>
</sourceDesc>
</fileDesc>
<profileDesc>
<textClass>
<keywords scheme="mix" xml:lang="en">
<term>Continuous Density Hidden Markov Models (CDHMMs)</term>
<term>Continuous Speech Recognition</term>
<term>Gaussian Mixture Models (GMMs)</term>
<term>Gaussian Mixture Models (GMMs).</term>
<term>Real Time Automatic Speech Recognition (ASR) System</term>
<term>Simplified Statistical Trainning Algorithm</term>
<term>Viterbi</term>
</keywords>
</textClass>
</profileDesc>
</teiHeader>
<front>
<div type="abstract" xml:lang="en">In this paper we present our own real time speaker-independent continuous phone recognition (Spirit) using Context-Independent Continuous Density HMMs (CI-CDHMMs) modeled by Gaussian Mixtures Models (GMMs). All the parameters of our system are estimated directly from data by using an improved Viterbi alignment process instead of the classical Baum-Welch estimation procedure. Generally, in the literature the Viterbi training algorithm is used as a pretreatment to initialize HMMs models that will be most often re-estimated by using complex re-estimation formula. In order to evaluate and compare the performance of our system with other previous works, we use the TIMIT database. The duration test of our recognition system for each sentence is between 2 seconds (for short sentences) to 12 seconds (for long sentences). We get, by combining the 64 possible phones into 39 phonetic classes, a phone recognition correct rate of 71.06% and an accuracy rate of 65.25%. These results compare favorably with previously published works.</div>
</front>
</TEI>
<affiliations>
<list>
<country>
<li>France</li>
<li>Maroc</li>
</country>
<region>
<li>Grand Est</li>
<li>Lorraine (région)</li>
</region>
<settlement>
<li>Nancy</li>
</settlement>
<orgName>
<li>Institut national polytechnique de Lorraine</li>
<li>Université Nancy 2</li>
<li>Université de Lorraine</li>
</orgName>
</list>
<tree>
<country name="Maroc">
<noRegion>
<name sortKey="Lachhab, Othman" sort="Lachhab, Othman" uniqKey="Lachhab O" first="Othman" last="Lachhab">Othman Lachhab</name>
</noRegion>
<name sortKey="Ibn Elhaj, El Hassan" sort="Ibn Elhaj, El Hassan" uniqKey="Ibn Elhaj E" first="El Hassan" last="Ibn Elhaj">El Hassan Ibn Elhaj</name>
</country>
<country name="France">
<region name="Grand Est">
<name sortKey="Di Martino, Joseph" sort="Di Martino, Joseph" uniqKey="Di Martino J" first="Joseph" last="Di Martino">Joseph Di Martino</name>
</region>
<name sortKey="Hammouch, Ahmed" sort="Hammouch, Ahmed" uniqKey="Hammouch A" first="Ahmed" last="Hammouch">Ahmed Hammouch</name>
</country>
</tree>
</affiliations>
</record>

Pour manipuler ce document sous Unix (Dilib)

EXPLOR_STEP=$WICRI_ROOT/Wicri/Lorraine/explor/InforLorV4/Data/Main/Exploration
HfdSelect -h $EXPLOR_STEP/biblio.hfd -nk 001B49 | SxmlIndent | more

Ou

HfdSelect -h $EXPLOR_AREA/Data/Main/Exploration/biblio.hfd -nk 001B49 | SxmlIndent | more

Pour mettre un lien sur cette page dans le réseau Wicri

{{Explor lien
   |wiki=    Wicri/Lorraine
   |area=    InforLorV4
   |flux=    Main
   |étape=   Exploration
   |type=    RBID
   |clé=     Hal:hal-00761816
   |texte=   Real Time Context-Independent Phone Recognition Using a Simplified Statistical Training Algorithm
}}

Wicri

This area was generated with Dilib version V0.6.33.
Data generation: Mon Jun 10 21:56:28 2019. Site generation: Fri Feb 25 15:29:27 2022